AITopics | information-theoretic generalization

We provide novel information-theoretic generalization bounds for stochastic gradient Langevin dynamics (SGLD) under the assumptions of smoothness and dissipativity, which are widely used in sampling and non-convex optimization studies. Our bounds are time-independent and decay to zero as the sample size increases, regardless of the number of iterations and whether the step size is fixed. Unlike previous studies, we derive the generalization error bounds by focusing on the time evolution of the Kullback-Leibler divergence, which is related to the stability of datasets and is the upper bound of the mutual information between output parameters and an input dataset. Additionally, we establish the first information-theoretic generalization bound when the training and test loss are the same by showing that a loss function of SGLD is sub-exponential. This bound is also time-independent and removes the problematic step size dependence in existing work, leading to an improved excess risk bound by combining our analysis with the existing non-convex optimization error bounds.

artificial intelligence, generalization, machine learning, (14 more...)

Neural Information Processing Systems

Country: Asia (0.28)

Technology:

Information Technology > Artificial Intelligence > Representation & Reasoning (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning > Gradient Descent (0.36)

Add feedback

fa67d13ba6c73637593bbcc92f6400ff-Paper-Conference.pdf

Neural Information Processing SystemsFeb-18-2026, 02:02:05 GMT

gen, generalization, generalization error, (14 more...)

Neural Information Processing Systems

Country:

North America > United States > Illinois > Champaign County > Urbana (0.14)
Europe > United Kingdom > England > Cambridgeshire > Cambridge (0.04)
North America > United States > Rhode Island > Providence County > Providence (0.04)
(2 more...)

Technology:

Information Technology > Artificial Intelligence > Representation & Reasoning (0.68)
Information Technology > Artificial Intelligence > Machine Learning > Learning Graphical Models (0.46)

Add feedback

Information Theoretic Lower Bounds for Information Theoretic Upper Bounds Roi Livni Department Electrical Engineering Tel Aviv University rlivni@tauex.tau.ac.il

Neural Information Processing SystemsFeb-14-2026, 20:16:25 GMT

But nonetheless, they tend to perform well on unseen test data.

artificial intelligence, information, machine learning, (16 more...)

Neural Information Processing Systems

Country:

Asia > Middle East > Israel > Tel Aviv District > Tel Aviv (0.40)
Europe > United Kingdom > England > Cambridgeshire > Cambridge (0.04)
North America > United States > California > Santa Clara County > Palo Alto (0.04)

Genre: Research Report (0.46)

Technology:

Information Technology > Artificial Intelligence > Machine Learning > Computational Learning Theory (0.69)
Information Technology > Artificial Intelligence > Representation & Reasoning (0.68)
Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning (0.47)

Add feedback

768396006e9214568dba5aae9dd312c5-Paper-Conference.pdf

Neural Information Processing SystemsFeb-14-2026, 20:16:22 GMT

algorithm, artificial intelligence, machine learning, (16 more...)

Neural Information Processing Systems

Country:

Europe > United Kingdom > England > Cambridgeshire > Cambridge (0.04)
North America > United States > California > Santa Clara County > Palo Alto (0.04)
Asia > Middle East > Israel > Tel Aviv District > Tel Aviv (0.04)

Genre: Research Report (0.46)

Technology:

Information Technology > Artificial Intelligence > Machine Learning > Computational Learning Theory (0.69)
Information Technology > Artificial Intelligence > Representation & Reasoning (0.68)
Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning (0.48)

Add feedback

5d632bbfde3f580d8183dc48ed87b418-Paper-Conference.pdf

Neural Information Processing SystemsFeb-14-2026, 17:23:38 GMT

artificial intelligence, generalization, machine learning, (17 more...)

Neural Information Processing Systems

Country:

North America > Canada > Ontario > Toronto (0.14)
Europe > United Kingdom > England > Cambridgeshire > Cambridge (0.04)
North America > Canada > Ontario > National Capital Region > Ottawa (0.04)
(2 more...)

Genre: Research Report > Experimental Study (0.93)

Industry: Information Technology (0.46)

Technology:

Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning (0.68)
Information Technology > Artificial Intelligence > Representation & Reasoning (0.68)

Add feedback

cf0d02ec99e61a64137b8a2c3b03e030-Paper.pdf

Neural Information Processing SystemsFeb-11-2026, 06:08:13 GMT

algorithm, generalization, information, (13 more...)

Neural Information Processing Systems

Country:

North America > United States > New York > New York County > New York City (0.04)
North America > United States > Illinois > Champaign County > Urbana (0.04)
Europe > United Kingdom > England > Cambridgeshire > Cambridge (0.04)

Industry:

Government (0.46)
Information Technology > Security & Privacy (0.46)

Technology:

Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (0.69)

Add feedback

41b6674c28a9b93ec8d22a53ca25bc3b-Paper-Conference.pdf

Neural Information Processing SystemsFeb-8-2026, 13:55:22 GMT

generalization, information, mutual information, (16 more...)

Neural Information Processing Systems

Country:

North America > Canada > Quebec > Montreal (0.04)
Europe > United Kingdom > England > Cambridgeshire > Cambridge (0.04)
Oceania > Australia > Victoria > Melbourne (0.04)
(11 more...)

Genre: Research Report (0.93)

Technology:

Information Technology > Artificial Intelligence > Representation & Reasoning (0.68)
Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning (0.46)

Add feedback

Time-Independent Information-Theoretic Generalization Bounds for SGLD

Neural Information Processing SystemsFeb-8-2026, 11:15:37 GMT

We provide novel information-theoretic generalization bounds for stochastic gradient Langevin dynamics (SGLD) under the assumptions of smoothness and dissi-pativity, which are widely used in sampling and non-convex optimization studies.

artificial intelligence, generalization, machine learning, (18 more...)

Neural Information Processing Systems

Country:

Europe > United Kingdom > England > Cambridgeshire > Cambridge (0.04)
Asia > Japan > Honshū > Kansai > Osaka Prefecture > Osaka (0.04)
Asia > Afghanistan > Parwan Province > Charikar (0.04)

Technology:

Information Technology > Artificial Intelligence > Representation & Reasoning (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.46)
Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning > Gradient Descent (0.35)

Add feedback

Filters

Collaborating Authors

information-theoretic generalization

Information about AI from the News, Publications, and Conferences

Automatic Classification – Tagging and Summarization – Customizable Filtering and Analysis

fa67d13ba6c73637593bbcc92f6400ff-Paper-Conference.pdf

19dbb86f771ddbf9986cf0c9b1c61c17-Supplemental-Conference.pdf

Time-Independent Information-Theoretic Generalization Bounds for SGLD

fa67d13ba6c73637593bbcc92f6400ff-Paper-Conference.pdf

Information Theoretic Lower Bounds for Information Theoretic Upper Bounds Roi Livni Department Electrical Engineering Tel Aviv University rlivni@tauex.tau.ac.il

768396006e9214568dba5aae9dd312c5-Paper-Conference.pdf

5d632bbfde3f580d8183dc48ed87b418-Paper-Conference.pdf

cf0d02ec99e61a64137b8a2c3b03e030-Paper.pdf

41b6674c28a9b93ec8d22a53ca25bc3b-Paper-Conference.pdf

Time-Independent Information-Theoretic Generalization Bounds for SGLD